The problem
10xCRM integrates more than a million records a day from its tenants, through Mage.ai pipelines on Google Cloud. Keeping those pipelines healthy cost 40 hours of paid maintenance a month.
Every failure was investigated by hand across Slack alerts, Mage.ai logs and the code in GitHub. The information existed, but it took 6–8 hours to piece it together and ship a fix.
How an incident is handled
What we built
Outcome
30 min
Issue resolution, down from six to eight hours
70%
Less manual operations work across the measured workflow
1M+
Tenant records integrated per day, across monitored pipelines
40 hrs
Monthly pipeline maintenance, now handled by the agent
For each failure, the agent proposes a known runbook, a configuration change, a replay or a code fix, with a complete diagnosis summary. Nothing reaches production on its own: an engineer reviews and merges every pull request. The 40 hours a month of maintenance 10xCRM used to pay for is now handled by the agent.
TECHNOLOGY
Google Cloud
Google BigQuery
Mage.ai
- Cloud Run
dbt
SlackGitHub
AWS
Amazon BedrockAnthropic Claude


